The LF Language Recognition System for NIST LRE 2011

نویسنده

  • Alberto Abad
چکیده

This document presents a description of INESC-ID’s Spoken Language Systems Laboratory (LF) Language Recognition systems submitted to the 2011 NIST Language Recognition evaluation. The LF primary system consists of the fusion of six individual sub-systems: four phonotactic sub-systems and two acoustic based sub-systems. The major differences of the submitted LR system with respect to previous LF system submitted to the NIST LRE 2009 campaign are: a) use of SVM discriminative modelling of expected phone counts extracted from lattices in contrast to generative n-gram modelling of phoneme sequences in phonotactic systems, b) development of a single kernel based system of Gaussian supervectors with support vector machine modelling, and c) incorporation of a new i-vector based system with linear generative classifiers. Additionally, two contrastive systems have been submitted. One fundamental particularity of the LF submission is that a relatively small training data set was defined and used for building the several sub-systems. Thus, the “small” training data set permitted fast development and comparison of algorithms and new subsystems.

برای دانلود رایگان متن کامل این مقاله و بیش از 32 میلیون مقاله دیگر ابتدا ثبت نام کنید

ثبت نام

اگر عضو سایت هستید لطفا وارد حساب کاربری خود شوید

منابع مشابه

The L2F Language Verification System for NIST LRE 2009

This paper presents a description of the INESC-ID’s Spoken Language Systems Laboratory (LF) Language Verification system submitted to the 2009 NIST Language Recognition evaluation. The LF system is composed by the fusion of eight individual sub-systems: four phonotactic systems and four acoustic based methods. Language recognition results have been submitted for the “closed-set”, “open-set” and...

متن کامل

The MITLL NIST LRE 2011 language recognition system

 This paper presents a description of the MIT Lincoln Laboratory (MITLL) language recognition system developed for the NIST 2011 Language Recognition Evaluation (LRE). The submitted system consisted of a fusion of four core classifiers, three based on spectral similarity and one based on tokenization. Additional system improvements were achieved following the submission deadline. In a major de...

متن کامل

Data selection and calibration issues in automatic language recognition - investigation with BUT-AGNITIO NIST LRE 2009 system

This paper summarizes the BUT-AGNITIO system for NIST Language Recognition Evaluation 2009. The post-evaluation analysis aimed mainly at improving the quality of the data (fixing language label problems and detecting overlapping speakers in the training and development sets) and investigation of different compositions of the development set. The paper further investigates into JFA-based acousti...

متن کامل

IIR System Description for the 2011 NIST Language Recognition Evaluation

The Institute for Infocomm Research (IIR) team submitted two systems, namely the primary iir_primary_llr and the contrastive iir_contrast1_llr, to the 2011 NIST Language Recognition Evaluation (LRE). Both systems are based on the fusion of multiple classifiers. These classifiers are broadly divided into two groups: acoustic and phonotactic. Included in the submission are the result files: iir1/...

متن کامل

Fusing language information from diverse data sources for phonotactic language recognition

The baseline approach in building phonotactic language recognition systems is to characterize each language by a single phonotactic model generated from all the available languagespecific training data. When several data sources are available for a given target language, system performance can be improved using language source-dependent phonotactic models. In this case, the common practice is t...

متن کامل

ذخیره در منابع من


  با ذخیره ی این منبع در منابع من، دسترسی به آن را برای استفاده های بعدی آسان تر کنید

عنوان ژورنال:

دوره   شماره 

صفحات  -

تاریخ انتشار 2011